AITopics | faster r-cnn

Collaborating Authors

faster r-cnn

Information about AI from the News, Publications, and Conferences

Automatic Classification – Tagging and Summarization – Customizable Filtering and Analysis

If you are looking for an answer to the question What is Artificial Intelligence? and you only have a minute, then here's the definition the Association for the Advancement of Artificial Intelligence offers on its home page: "the scientific understanding of the mechanisms underlying thought and intelligent behavior and their embodiment in machines."

However, if you are fortunate enough to have more than a minute, then please get ready to embark upon an exciting journey exploring AI (but beware, it could last a lifetime) …

PrObeD: Proactive Object Detection Wrapper

Neural Information Processing SystemsFeb-18-2026, 00:02:00 GMT

These works are regarded as passive works for object detection as they take the input image as is. However, convergence to global minima is not guaranteed to be optimal in neural networks; therefore, we argue that the trained weights in the object detector are not optimal. To rectify this problem, we propose a wrapper based on proactive schemes, PrObeD, which enhances the performance of these object detectors by learning a signal. PrObeD consists of an encoder-decoder architecture, where the encoder network generates an image-dependent signal termed templates to encrypt the input images, and the decoder recovers this template from the encrypted images. We propose that learning the optimum template results in an object detector with an improved detection performance. The template acts as a mask to the input images to highlight semantics useful for the object detector. Finetuning the object detector with these encrypted images enhances the detection performance for both generic and camouflaged.

artificial intelligence, detector, machine learning, (18 more...)

Neural Information Processing Systems

Country:

South America > Brazil (0.04)
North America > United States > Michigan (0.04)
Asia (0.04)

Industry:

Information Technology (0.47)
Health & Medicine (0.46)

Technology:

Information Technology > Sensing and Signal Processing > Image Processing (1.00)
Information Technology > Artificial Intelligence > Vision (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.69)

Add feedback

RelationNet++: Bridging Visual Representations for Object Detection via Transformer Decoder

Neural Information Processing SystemsDec-24-2025, 08:58:16 GMT

Existing object detection frameworks are usually built on a single format of object/part representation, i.e., anchor/proposal rectangle boxes in RetinaNet and Faster R-CNN, center points in FCOS and RepPoints, and corner points in CornerNet. While these different representations usually drive the frameworks to perform well in different aspects, e.g., better classification or finer localization, it is in general difficult to combine these representations in a single framework to make good use of each strength, due to the heterogeneous or non-grid feature extraction by different representations.

artificial intelligence, proceedings, representation, (7 more...)

Neural Information Processing Systems

Technology:

Information Technology > Artificial Intelligence (0.60)
Information Technology > Data Science (0.41)

Add feedback

A Physics-Constrained, Design-Driven Methodology for Defect Dataset Generation in Optical Lithography

Hu, Yuehua, Kong, Jiyeong, Shin, Dong-yeol, Kim, Jaekyun, Kang, Kyung-Tae

arXiv.org Artificial IntelligenceDec-11-2025

The efficacy of Artificial Intelligence (AI) in micro/nano manufacturing is fundamentally constrained by the scarcity of high-quality and physically grounded training data for defect inspection. Lithography defect data from semiconductor industry are rarely accessible for research use, resulting in a shortage of publicly available datasets. To address this bottleneck in lithography, this study proposes a novel methodology for generating large-scale, physically valid defect datasets with pixel-level annotations. The framework begins with the ab initio synthesis of defect layouts using controllable, physics-constrained mathematical morphology operations (erosion and dilation) applied to the original design-level layout. These synthesized layouts, together with their defect-free counterparts, are fabricated into physical samples via high-fidelity digital micromirror device (DMD)-based lithography. Optical micrographs of the synthesized defect samples and their defect-free references are then compared to create consistent defect delineation annotations. Using this methodology, we constructed a comprehensive dataset of 3,530 Optical micrographs containing 13,365 annotated defect instances including four classes: bridge, burr, pinch, and contamination. Each defect instance is annotated with a pixel-accurate segmentation mask, preserving full contour and geometry. The segmentation-based Mask R-CNN achieves AP@0.5 of 0.980, 0.965, and 0.971, compared with 0.740, 0.719, and 0.717 for Faster R-CNN on bridge, burr, and pinch classes, representing a mean AP@0.5 improvement of approximately 34%. For the contamination class, Mask R-CNN achieves an AP@0.5 roughly 42% higher than Faster R-CNN. These consistent gains demonstrate that our proposed methodology to generate defect datasets with pixel-level annotations is feasible for robust AI-based Measurement/Inspection (MI) in semiconductor fabrication.

artificial intelligence, defect, machine learning, (18 more...)

arXiv.org Artificial Intelligence

2512.09001

Country:

Asia > South Korea > Seoul > Seoul (0.04)
Asia > Myanmar > Tanintharyi Region > Dawei (0.04)

Genre: Research Report (1.00)

Industry:

Semiconductors & Electronics (1.00)
Information Technology > Hardware (0.48)

Technology:

Information Technology > Artificial Intelligence > Vision (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Statistical Learning (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks > Deep Learning (0.69)

Add feedback

R-FCN: Object Detection via Region-based Fully Convolutional Networks

jifeng dai, Yi Li, Kaiming He, Jian Sun

Neural Information Processing SystemsNov-21-2025, 06:27:55 GMT

VOC datasets (e.g., 83.6% mAP on the 2007 set) with the 101-layer ResNet. Meanwhile, our result is achieved at a test-time speed of 170ms per image, 2.5-20

artificial intelligence, convolutional layer, machine learning, (17 more...)

Neural Information Processing Systems

Country:

Europe > Spain > Catalonia > Barcelona Province > Barcelona (0.04)
Asia (0.04)

Technology:

Information Technology > Artificial Intelligence > Vision (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (1.00)

Add feedback

acaa23f71f963e96c8847585e71352d6-Paper.pdf

Neural Information Processing SystemsNov-15-2025, 01:08:34 GMT

computer vision, machine learning, natural language, (17 more...)

Neural Information Processing Systems

Country:

North America > United States > Hawaii > Honolulu County > Honolulu (0.04)
North America > United States > Massachusetts > Suffolk County > Boston (0.04)
North America > United States > California > Los Angeles County > Long Beach (0.04)
(2 more...)

Technology:

Information Technology > Artificial Intelligence > Vision (1.00)
Information Technology > Artificial Intelligence > Representation & Reasoning (1.00)
Information Technology > Artificial Intelligence > Natural Language (1.00)
Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.68)

Add feedback

acaa23f71f963e96c8847585e71352d6-AuthorFeedback.pdf

Neural Information Processing SystemsNov-15-2025, 01:08:22 GMT

artificial intelligence, cobe, machine learning, (15 more...)

Neural Information Processing Systems

Technology: Information Technology > Artificial Intelligence > Machine Learning (0.50)

Add feedback

REVIVE: Regional Visual Representation Matters in Knowledge-Based Visual Question Answering (Supplementary Materials) A Overview

Neural Information Processing SystemsNov-14-2025, 03:14:01 GMT

In the supplementary materials, we provide the following sections: (a) Implementation details of implicit knowledge retrieval in Section B. (b) Ablation study experiments in Section C. (c) Visualization results in Section D. We first describe more implementation details of implicit knowledge retrieval of the proposed REVIVE . Specifically, we explain how we extract multiple answer candidates. PICa's multi-query ensemble approach, we take all these In our experiments, we just retrieve 5 ( i.e., When using only one implicit knowledge candidate, the model can achieve 55.8% accuracy, However, when the retrieved candidate number is 8, we can see that the performance isn't the best, Table 3: Ablation study on using different object detectors. R-CNN (R101) mean using ResNet-50 [2] and ResNet-101 [2] as backbones. We can see that Faster R-CNN with ResNet-50 and ResNet-101 as the backbone can achieve 55.3% and 55.6% accuracy respectively, and using the GLIP as the object detector can achieve the optimal Conference on Computer Vision and Pattern Recognition, pages 3195-3204, 2019. 2 Figure 1: The implicit knowledge retrieval visualization results without and with the proposed regional descriptions/tags.

knowledge management, machine learning, pattern recognition, (17 more...)

Neural Information Processing Systems

Genre: Research Report > New Finding (0.35)

Technology: